Tag
1 article
The FineBooks project from Hugging Face and EleutherAI tested 14 open-source OCR models on more than 2,000 historical book pages, with the top model achieving 97.6% character accuracy. While this is sufficient for AI training data, it falls short of scholarly transcription standards.